Skip to content

support transformers 5.14.1 - #2024

Merged
hhaAndroid merged 2 commits into
InternLM:mainfrom
hhaAndroid:support_14
Aug 18, 2026
Merged

hhaAndroid merged 2 commits into
InternLM:mainfrom
hhaAndroid:support_14

Conversation

@hhaAndroid

@hhaAndroid hhaAndroid commented Aug 14, 2026 •

Copy link
Copy Markdown
Collaborator
  • 修复升级到 transformer 5.14.1 带来的 video processor 处理逻辑不匹配问题(HF video processor 有 BC)
  • 修复升级到 transformer 5.14.1 带来的 qwen3.5 ci 数值精度不对齐问题(HF topk weight 精度有 BC)

@hhaAndroid
hhaAndroid merged commit 58710bc into InternLM:main Aug 18, 2026
8 checks passed
longboat2010 added a commit to Ascend-SHL-SACT/xtuner that referenced this pull request Aug 20, 2026
…Split) (#30)

* [ci] change lmdeploy version and related packages (InternLM#1996)

* [ci] update vllm case in gpu

* fir error on rl vllm case

* fix script error

* fix error

* [Fix] Fix Ray generate concurrency group metadata (InternLM#2005)

* Fix Ray generate concurrency group metadata

* fix docs build error

* [Fix] Avoid scanning Transformers lazy module during test collection (InternLM#2008)

* test: avoid scanning Transformers lazy module

* test: call dense decoder with keyword arguments

* test: preserve Qwen3.5 vision interpolation dtype

* Fix custom preprocess and postprocess in judger pools (InternLM#2000)

* [ci] Add Intern-S2-Preview RL ETE coverage and user examples (InternLM#2010)

* [CI] Expand Claude review guidance (InternLM#2012)

* [CI] Use latest Claude Opus model (InternLM#2013)

* [Fix] discard expired states when tail batch is disabled (InternLM#2006)

* discard expired states when tail batch is disabled

* delete bind function and add retryable attr

* [GLM-5.2] Preserve HF config compatibility for vLLM (InternLM#1998)

* [GLM-5.2] Preserve HF export compatibility fields

Keep legacy routing metadata required by vLLM and trim unused RoPE defaults from exported configs.

* [Testing] Add HF config export contract checker

* [Skills] Handle missing HF config exporters

* [Fix] Supervise GLM-5.2 assistant stop tokens (InternLM#1997)

* [Fix] Supervise GLM-5.2 assistant stop tokens

* [SKILL] Add chat-template audit and implementation skill

Co-authored-by: Cursor <cursoragent@cursor.com>

* [SKILL] Remove Python environment instruction

---------

Co-authored-by: Cursor <cursoragent@cursor.com>

* [CI] Improve Claude action completion and tool access (InternLM#2015)

* [CI] Improve Claude review completion

* [CI] Expose Claude review turn budget

* [CI] Expand sandboxed Claude review tools

* [CI] Rely on isolated review runner

* [CI] Expand Claude comment capabilities

* [CI] Define Claude review section order

* [CI] Add Claude test review guidance

* [Fix] Add GLM-5.2 MuonSplit and AdamW-only gradient clipping (InternLM#2001)

* Fix GLM-5.2 Muon splitting and gradient clipping

* Fix MuonSplit parameter typing

* Refactor gradient clipping policy

* [CI] Split Claude review analysis and publication (InternLM#2016)

* [CI] Split Claude review into analysis and publish phases

* fix(ci): publish Claude review summary deterministically

* fix(ci): scope Claude review to prepared PR diff

* fix(ci): harden Claude review credentials

* fix(ci): prepare deterministic Claude review bundle

* refactor(ci): simplify Claude review workflow

* [CI] Streamline Claude review output and publication (InternLM#2019)

* [CI] Fix Claude review summary publication

Allow the StructuredOutput tool required by --json-schema instead of denying every tool, and give the summary phase a second turn. Limit reported findings to Warning severity and above to keep reviews concise.

* [CI] Streamline Claude review prompt and deduplication

Limit review output to supported Warning+ findings, preserve the summary contract, and add a normalized discussion index for efficient deduplication.

* [CI] Refine Claude review comment format

Classify inline findings, keep routine comments within 150 characters, and state the required summary section order.

* add qwen3.5 test cases about ep_size (InternLM#2011)

* add ep case

* update config

* add rl mtp config

* debug

* update step

* debug

* debug

* add ci debug env

* update config

* update description

* [ci] optimize ete false positive (InternLM#2028)

* Reduce ETE false positives from tight KL/time thresholds and offline resume first checks.

Loosen qwen3-5 VL RL mismatch_k3_kl and qwen3-rl-lmdeploy KL/time budgets, and compare only the first-run tracker prefix when phase=first sees a merged offline file.

* Check mismatch_k3_kl per-step value < 0.001 instead of baseline drift.

Add method=value for RL metric bounds and apply it to all K3 KL checks so ETE no longer fails on run-to-run abs diffs.

* Align qwen3-5-rl-vl-lmdeploy-mtp-ep K3 KL check to per-step value < 0.001.

* update

* Harden ETE checks: fix RL resume meta, drop SFT round(,2), slowdown-only time.

Resolve colocate RL meta for update_meta, compare SFT relative error without two-decimal rounding, and only penalize time/step regressions for s2-preview.

* drop faild (InternLM#1945)

Improve agent rollout failure handling

* [Feat] Add off-policy masking for partial rollouts (InternLM#2003)

* Add token staleness masking during batch take

* Add lifecycle-aware sampling for token-expired groups

* Refine expired rollout sampling and tail-batch semantics

* Refine staleness lifecycle for expired groups

* fix claude comments

* [Fix] fix  RL MTP config handling for non-compose models (InternLM#2031)

Fix RL MTP config handling for non-compose models

* support transformers 5.14.1 (InternLM#2024)

* support transformers 5.14.1

* update

* [ci] Optimize ete case, relax sft metric after dropping round (InternLM#2034)

* [ci] relax SFT metric thresholds after dropping round(,2)

Align loss and text_tokens tolerances with observed run-to-run drift so ETE baseline checks match unrounded relative error semantics.

Co-authored-by: Cursor <cursoragent@cursor.com>

* [ci] use absolute diff for text_tokens in SFT metric checks

Add per-metric comparison method support and compare token counts with exact-match threshold 0 while keeping relative drift checks for loss metrics.

Co-authored-by: Cursor <cursoragent@cursor.com>

---------

Co-authored-by: JuliaLin <julialin@JuliaLindeMacBook-Pro.local>
Co-authored-by: Cursor <cursoragent@cursor.com>

* [Fix] Preserve concurrent trace sessions in disaggregated RL (InternLM#2021)

* [Fix] Preserve concurrent rollout trace sessions

* [Fix] Address trace-session lifecycle review

* refactor(rl): centralize terminal trace cleanup

* refactor(rl): use non-retryable cleanup terminology

* style(rl): apply docformatter

* fix(rl): invalidate trace store cache without Ray

---------

Co-authored-by: zhulinJulia24 <145004780+zhulinJulia24@users.noreply.github.com>
Co-authored-by: kkscilife <126147887+kkscilife@users.noreply.github.com>
Co-authored-by: Yanhui Duan <dyh10280@163.com>
Co-authored-by: Penghao Zhao <henryzhao1989@163.com>
Co-authored-by: Cursor <cursoragent@cursor.com>
Co-authored-by: 纪焘 (Tao Ji) <taoji.cs@gmail.com>
Co-authored-by: liukuikun <24622904+Harold-lkk@users.noreply.github.com>
Co-authored-by: PengchengShi00 <146822991+PengchengShi00@users.noreply.github.com>
Co-authored-by: Haian Huang(深度眸) <1286304229@qq.com>
Co-authored-by: JuliaLin <julialin@JuliaLindeMacBook-Pro.local>
Co-authored-by: matrix72 <60974665+matrix72c@users.noreply.github.com>
@hhaAndroid
hhaAndroid deleted the support_14 branch August 25, 2026 02:00
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

2 participants